Back

Journal of Clinical Pathology

BMJ

Preprints posted in the last 7 days, ranked by how well they match Journal of Clinical Pathology's content profile, based on 15 papers previously published here. The average preprint has a 0.02% match score for this journal, so anything above that is already an above-average fit.

1
Maternal cell-free RNA versus combined screening for first-trimester prediction of early-onset preeclampsia: a nested case-control study

Satorres-Perez, E.; Castillo-Marco, N.; Igual, M.; Cordero, T.; Munoz-Blat, I.; Monfort-Ortiz, R.; Marcos-Puig, B.; Simon, C.; Garrido-Gomez, T.; Perales-Marin, A.

2026-09-02 obstetrics and gynecology 10.64898/2026.08.28.26361628 medRxiv
Top 0.1%
7.9%
Show abstract

Background. In Europe, first-trimester combined screening with the Fetal Medicine Foundation (FMF) algorithm identifies women at increased risk of preeclampsia who may benefit from personalized aspirin prophylaxis. However, a substantial proportion of early-onset preeclampsia (EOPE) remains undetected at clinically acceptable specificity. Objective. To evaluate the first-trimester performance of MaiRa for early-onset preeclampsia (EOPE) risk stratification by benchmarking it against FMF screening in the same women, characterizing discordant patient-level classification profiles and exploring potential implementation strategies. Study Design. This secondary case-control analysis was nested within the prospective, multicentre PREMOM cohort [NCT04990141], which enrolled women with singleton pregnancies across 14 tertiary hospitals in Spain. First-trimester MaiRa and FMF risk estimates were evaluated in the same 126 pregnant women, comprising 99 uncomplicated controls and 27 EOPE cases, defined by disease onset before 34 weeks. Discrimination was compared using a stratified paired bootstrap analysis of the areas under the receiver-operating-characteristic curves. Performance was assessed at prespecified clinical thresholds, and detection rates were evaluated at fixed false-positive rates. Universal and contingent MaiRa implementation strategies were also evaluated. Results. MaiRa showed greater first-trimester discrimination for EOPE than FMF combined screening (AUC, 0.974 vs 0.900; P=.040) and consistently achieved higher detection rates across fixed false-positive rates. At false-positive rates of 5% and 10%, MaiRa detected 85.2% and 92.6% of EOPE cases, compared with 44.4% and 70.4% for FMF, respectively. Patient-level analysis demonstrated that MaiRa identified 12 of 27 EOPE cases (44.4%) classified as low risk by FMF; these pregnancies generally exhibited less abnormal conventional first-trimester profiles, including fewer maternal risk factors, lower mean arterial pressure and lower uterine artery pulsatility index, yet 8 of 12 (66.7%) subsequently developed severe EOPE. Exploratory implementation analyses showed that universal MaiRa screening achieved the highest EOPE detection, whereas a contingent strategy using FMF for triage and reflex MaiRa testing reduced molecular testing to 35.7% of pregnancies while maintaining 77.8% sensitivity and 97.0% specificity. Conclusion. MaiRa provided greater first-trimester discrimination for EOPE than conventional combined screening and detected additional pregnancies that later developed severe disease despite less abnormal conventional screening profiles. The findings suggest that maternal plasma cfRNA profiling captures biological alterations not fully reflected by combined first-trimester screening and support further prospective evaluation in an independent, unselected obstetric population. Key words: early-onset preeclampsia; first-trimester screening; cell-free RNA; liquid biopsy; Fetal Medicine Foundation algorithm; combined screening; risk stratification; aspirin prophylaxis.

2
Burden of fatigue in compensated chronic liver disease: findings from the multinational a:GAP Study

Choudhuri, G.; Akhundova-Unadkat, G.; Naidoo, N.; Morales-Castillo, M.; Guillaume, X.; Duijnhoven, R. G.; Safaei, A.; Swain, M. G.

2026-09-02 gastroenterology 10.64898/2026.08.28.26361618 medRxiv
Top 0.2%
1.6%
Show abstract

Background & Aims: Fatigue is a central symptom of chronic liver disease (CLD), substantially impacting health-related quality of life (HRQoL). This study aimed to further understand CLD symptomatology, including fatigue, and its impact on HRQoL from a patient perspective. Methods: Abbott Global Assessment of Patients unmet needs (aGAP) was a multinational, cross-sectional survey in adults with compensated CLD in China, India and Mexico, conducted between July and November 2024. Adult participants who self-reported that they had physician-diagnosed CLD and were experiencing fatigue completed a quantitative survey to assess symptom burden and included three HRQoL patient-reported outcome (PRO) questionnaires (Patient-Reported Outcomes Measurement Information System [PROMIS]-29+2, Work Productivity and Activity Impairment - Specific Health Problem version 2.0 [WPAI: SHP], Multidimensional Fatigue Inventory [MFI]). Results: Overall, 505 participants (China: 200; Mexico: 105; India: 200) completed the study. Participants reported that their CLD-related fatigue sometimes, often or always affected their self-esteem/confidence (45.1%) and ability to maintain or acquire new employment (38.6%). Most participants reported moderate (51.3%) or serious (26.9%) fatigue, with 33.5% experiencing fatigue every day or almost every day. Many participants felt their social life was negatively impacted by their fatigue (47.3%) and that there were related financial difficulties (53.9%). Use of validated PRO tools demonstrated severe fatigue (MFI: overall mean [SD] 13.9 [3.4] general fatigue and 13.4 [3.6] physical fatigue) as well as substantial levels of work and activity impairment (WPAI: SHP overall mean [SD] 53.0 [26.4]) and high levels of anxiety, pain interference, depression and sleep interference (PROMIS T-scores [≥]54). Conclusions: Fatigue has a substantial impact on HRQoL among adults with CLD across several countries, highlighting a global unmet need for targeted interventions to effectively identify and manage the condition.

3
Patient safety culture, teamwork, and observed perioperative safety compliance in a high-volume surgical unit

Sahputri, V.; Angeline, A.; Tenggono, E.

2026-09-02 health systems and quality improvement 10.64898/2026.08.31.26361796 medRxiv
Top 0.3%
1.1%
Show abstract

Perioperative safety checklists standardize critical actions, but reliable completion depends on the surrounding work system and team behavior. We conducted a prospective observational analytic study from April to May 2026 in the central surgical unit of a high-volume public teaching referral hospital in Indonesia to examine whether patient safety culture and teamwork were associated with directly observed perioperative safety compliance and whether teamwork mediated the culture-compliance relationship. Patient safety culture was measured with the Hospital Survey on Patient Safety Culture 2.0, teamwork with a 35-item TeamSTEPPS Teamwork Perceptions Questionnaire research adaptation, and compliance by direct role-based observation using a 45-item checklist derived from the AORN Comprehensive Surgical Checklist. Eighty of 92 recruited professionals contributed 240 person-operation observations across 50 operations. Overall compliance was 74.75%, with sign-out lowest at 70.68%. Patient safety culture was associated with teamwork ({beta} = 0.590; 95% CI 0.510-0.770) and directly with compliance ({beta} = 0.407; 95% CI 0.187-0.712). The teamwork-compliance coefficient was positive ({beta} = 0.285; p = 0.046), but the prespecified percentile 95% CI included zero (-0.045 to 0.517). The indirect effect through teamwork was not supported ({beta} = 0.168; p = 0.079). These findings support a system-level interpretation of perioperative safety and identify learning-oriented responses to error, situation monitoring, and sign-out fidelity as measurable targets for future improvement efforts.

4
Electronic health data exploring cardiorespiratory responses of transfusions in preterm infants: An international multicenter cohort study

Honore, A.; Rech, T.; Scrivens, A.; Binotto, I.; Zandvoort, C. S.; van der Staaij, H.; Peck, M.; Zivanovic, S.; Stanworth, S. J.; Hartley, C.; Dame, C.; Deschmann, E.; the Neonatal Transfusion Network,

2026-09-03 pediatrics 10.64898/2026.09.01.26361418 medRxiv
Top 0.4%
0.9%
Show abstract

Background and Objectives: Preterm infants are commonly transfused, yet direct cardiorespiratory effects of red blood cell (RBC) transfusions remain poorly understood. We explored the feasibility of using multicentre electronic health data (EHD) to study such cardiorespiratory responses. Methods: Highly granular routine EHD were collected from preterm infants born <32 weeks gestational age at three European centres. Heart rate, oxygen saturation, and respiratory rate were evaluated 12 hours before and after the RBC transfusion. Results: A total of 321 transfusions in 164 infants were analysed. Overall, there was no significant change in the rate of bradycardia and apnoea following transfusion. Cardiorespiratory parameters varied substantially between infants; e.g. 20% of transfusions were associated with an unexpected, significant increase in heart rate. Respiratory rate and oxygen saturation exhibited similarly heterogenous patterns following transfusion. In sub-group analysis, the proportion of transfusions with increased heart rate was significantly higher within the first two weeks than later (32% vs 13%, p=0.0019). Conclusions: Multicentre EHD extraction allows to identify otherwise masked short-term effects of RBC transfusions on cardiorespiratory parameters, possibly indicating cardiac or pulmonary overload. Such effects may vary with adaptation to anaemia. Analysing EHD may ultimately enable personalized transfusion practice.

5
Clinical features of COVID-19 patients hospitalized at the Tashkent State Medical University and risk factors for intensive care unit admission: a cross-sectional study from Uzbekistan, Central Asia

Rakhimov, B.; Choi, J.; Kim, K.; Tuychiev, L.; Shadmanov, A.; Mamatkulov, B.

2026-08-31 infectious diseases 10.64898/2026.08.28.26361631 medRxiv
Top 0.6%
0.5%
Show abstract

Background. The clinical course of coronavirus disease 2019 (COVID-19), and the ability to anticipate which patients will require intensive care, were poorly characterized in Central Asia during the first pandemic wave. We aimed to describe the clinical features of hospitalized COVID-19 patients at the Tashkent State Medical University, Uzbekistan, and to identify risk factors for intensive care unit (ICU) admission. Methods. In this single-centre cross-sectional study, we reviewed the records of 2500 consecutive patients hospitalized between 11 April and 8 August 2020. Patients were grouped as asymptomatic or symptomatic, and symptomatic patients were compared by ICU versus non-ICU status. Groups were compared with chi-square or Fisher's exact and Mann-Whitney U tests. Univariable and multivariable logistic regression identified risk factors for ICU admission. Results. Of 2500 patients (median age 36 years; 60.9% male), 989 (39.6%) were asymptomatic and 1511 (60.4%) symptomatic. In total, 129 (5.2%) were admitted to the ICU and 38 (1.5%) died. ICU patients were older (median 56 vs 40.5 years) and more often had bilateral pneumonia, oxygen desaturation and cardiometabolic comorbidity. In the multivariable model (AUC 0.82), the independent predictors of ICU admission were ischemic heart disease (aOR 4.20), shortness of breath (aOR 3.22), hypertensive heart disease (aOR 2.93) and male sex (aOR 2.00). Conclusions. Older age, cardiometabolic comorbidity and respiratory compromise identified patients at high ICU risk. As one of the first clinical COVID-19 descriptions from Uzbekistan, these data provide a baseline for preparedness in Central Asia.

6
IL-10 Overexpression Improves Cerebral Microcirculation and Attenuates Cerebral Vasospasm After Experimental SAH

Nogami, K.; Ishii, H.; Demura, M.; Nakamura, T.; Loc, N. D.; Takarada-Iemata, M.; Tsunekawa, Y.; Nitahara-Kasahara, Y.; Okada, T.; Kamide, T.; Nakada, M.; Hori, O.

2026-08-29 pathology 10.64898/2026.08.25.747167 medRxiv
Top 0.6%
0.5%
Show abstract

BACKGROUND: Subarachnoid hemorrhage (SAH) induces inflammatory responses and subsequent immune cell activation, which may contribute in cerebral vasospasm, microcirculatory impairment and poor neurological outcomes. Although cerebral vasospasm has traditionally been considered a major cause of delayed cerebral ischemia after SAH, therapies targeting angiographic vasospasm have not consistently improved functional outcomes. Early inflammatory responses may contribute to microcirculatory impairment, cerebral vasospasm, and subsequent neurological injury. Herein, we investigated whether interleukin-10 (IL-10), an anti-inflammatory cytokine, improves these outcomes in an experimental SAH model. METHODS: Mice received intramuscular injections of either an adeno-associated virus encoding IL-10 (AAV/IL-10) vector or an AAV expressing green fluorescent protein (AAV/GFP) vector (control). India ink angiography was performed to assess the diameter of the sphenoidal segment of the middle cerebral artery (MCA), the total length of the visible cortical arteries, and cortical staining intensity, as indices of cerebral vasospasm, microcirculatory impairment, and cerebral perfusion, respectively. Perivascular inflammatory cell infiltration and cytokine levels were assessed using immunohistochemistry and ELISA. We also evaluated the therapeutic efficacy of the AAV/IL-10 vector when administered immediately after SAH induction. RESULTS: IL-10 overexpression significantly improved neurological outcomes after SAH and was associated with attenuated cerebral vasospasm and microcirculatory impairment, as well as preservation of cerebral perfusion. It also significantly reduced neutrophil and macrophage infiltration around the internal carotid artery and attenuated SAH-induced elevations in IL-6 and matrix metalloproteinase-3 levels. Mice treated with the AAV/IL-10 vector immediately after SAH induction showed significant improvements in neurological scores and cerebral perfusion. CONCLUSIONS: AAV-mediated IL-10 overexpression improves neurological outcomes after SAH, likely by attenuating inflammatory responses, cerebral vasospasm, and microcirculatory impairment. These findings suggest that IL-10-based anti-inflammatory therapy is a promising therapeutic strategy for SAH.

7
Exploring the negative triad of childhood maltreatment, fear of relapse, and low sleep quality in multiple sclerosis

Karabatsiakis, A.; Trepel, N.; Gander, M.; Buchheim, A.

2026-09-03 health systems and quality improvement 10.64898/2026.08.31.26361813 medRxiv
Top 0.7%
0.4%
Show abstract

Background: Multiple sclerosis (MS) is a chronic, immune-mediated disease of the central nervous system marked by demyelination and neurodegeneration. Beyond physical symptoms, MS is often linked to clinically relevant sleep disturbances. The variability and unpredictability of symptoms and disease progression can also fuel fear of relapse (FoR), undermining well-being and potentially increasing morbidity through inflammatory processes. Understanding biopsychosocial risk factors, including childhood maltreatment (CM) and sleep, in relation to FoR remains an important gap in MS management and research. Methods: Data from N = 48 participants were collected via an online survey. We used the Pittsburgh Sleep Quality Index (PSQI), the Fear-of-Relapse Scale (FoR), and the Childhood Trauma Questionnaire (CTQ) to assess the variables of interest. In addition, time points of exposure to different CM subtypes were assessed. Linear regression analyses were conducted to examine associations within the proposed negative triad. Results: A significant negative association between overall sleep quality and FoR was observed. In the total cohort, the interaction between CM and sleep was not a significant predictor of FoR. However, exploratory analysis revealed a significant interaction between CM and sleep among male participants, whereas the same interaction was not significant among female participants. Conclusion: A history of CM and impaired sleep quality introduce new stressors in managing one's own illness that have received little attention to date. However, the present study found that these factors were at least partly influential on the FoR. The results underscore the translational need for additional support services to enhance prevention and personalized care.

8
Increasing Lung Cancer Screening Participation Using an Informational Video Nudge: A Randomized Feasibility Trial

Wain, K. F.; Carroll, N. M.; Maclennan, A. J.; Hixon, B.; Steiner, J.; Ritzwoller, D. P.

2026-09-01 health systems and quality improvement 10.64898/2026.08.28.26361654 medRxiv
Top 0.7%
0.4%
Show abstract

Purpose: Lung cancer screening (LCS) with low-dose computed tomography (LDCT) reduces lung cancer mortality, yet screening participation remains low. We evaluated whether a brief informational video nudge delivered immediately before a scheduled clinical encounter increased LCS ordering and baseline LCS completion. Patients and Methods: We conducted a randomized feasibility trial within Kaiser Permanente Colorado from March through October 2025. LCS-eligible patients with an upcoming primary care or pulmonology appointment were assigned to intervention or usual care based on birth month. Intervention patients were split into two group, a group who received the LCS informational video nudge via text message within 24 hours of an eligible appointment; and second group who received the text plus a QR code video link during appointment rooming. Outcomes included LCS orders, baseline LCS-LDCT completion, and video engagement. Multivariable logistic regression was used to evaluate factors associated with LCS ordering. Results: Among 1,093 patients, 549 were assigned to intervention and 544 to usual care. Intervention patients were more likely to receive an LCS order within 1 day of their appointment (22.6% vs 16.4%; p=.010) and any time during follow-up (32.6% vs 24.1%; p=.002). Baseline LCS-LDCT completion was 51% higher in the intervention group, although the difference was not statistically significant (8.6% vs 5.7%; p=.078). Among the intervention group, 93 individuals (17%) viewed the video, generating 114 total views, and viewers watched an average of 79% of the video. Most views (82.5%) occurred through text-message delivery rather than QR codes. Conclusion: A brief, low-burden LCS informational video delivered immediately before a clinical encounter and integrated into existing workflows significantly increased LCS ordering and was associated with higher screening completion. Timely, scalable digital nudges may provide an effective strategy for improving LCS participation. Based on the observed effectiveness, feasibility, and efficiency of the intervention, KPCO incorporated the behavioral nudge into standard clinical care in February 2026.

9
Hormonal Therapies For Endometriosis: A Systematic Review And Meta-Analysis Of Randomised Head-To-Head Trials

Bandini, V.; Whitaker, L. H.; Vincent, K.; Salmeri, N.; Mawson, R.; Vercellini, P.; Horne, A. W.

2026-08-31 obstetrics and gynecology 10.64898/2026.08.26.26361442 medRxiv
Top 0.8%
0.4%
Show abstract

Background: Endometriosis is a chronic pain condition in which hormonal therapies form the cornerstone of long-term management. Treatment tolerability is critical for adherence and therapeutic success, but most comparative studies and reviews have focused on their ability to reduce menstrual pain, while their impact on non-menstrual pelvic pain (NMPP), bleeding patterns, adverse events (AEs), treatment discontinuation and quality of life (QoL) remain poorly characterised. This systematic review and meta-analysis evaluate these outcomes across currently available hormonal therapies, providing practical evidence for clinical decision-making. Methods: PubMed/MEDLINE, Scopus, and Embase were searched up to November 2025 for randomised controlled trials comparing at least two active first- or second-line hormonal treatments for endometriosis. Studies without confirmed endometriosis, treatment duration less than three months and comparing therapies to placebo only were excluded. Data were extracted by two reviewers from reports. Pain outcomes were pooled as mean differences (MD, 95% CI), with bleeding patterns, AEs, and discontinuations as proportions. Analyses were performed in R. PROSPERO: CRD420251137785. Findings: Of 1892 records screened, 48 trials (5583 women) met our inclusion criteria. Overall pelvic pain (0-10 scale) was significantly reduced across all treatment categories (p<0.001): combined oral contraceptives (COCs) (MD 3.17), oral and long-acting progestogens (MD 3.83; MD 4.29), and GnRH-analogues (MD 3.81). Sensitivity analyses restricted to studies reporting NMPP yielded comparable results. GnRH-agonists showed the most favourable bleeding profile, followed by continuous COCs. However, all regimens reported class-specific AEs, including mood changes, nausea, headache, weight gain, and decreased libido (pooled proportions >10%). Overall discontinuation due to AEs was 7.7%, and vaginal bleeding was the leading cause. Heterogeneity across meta-analyses was high. Risk of bias (RoB2) was moderate to high. Interpretation: Given similar reductions in overall pelvic pain across hormonal therapies, treatment decisions should prioritise differences in bleeding profiles, therapy-specific AEs, and QoL. Funding: None.

10
The effectiveness of a complex intervention, aimed at reducing hospital occupancy, to improve Emergency Department patient flow: a retrospective controlled interrupted time series

McHenry, R. D.; Caesar, D.; Clarke, B.; Mackay, D.; Pell, J.

2026-09-03 health systems and quality improvement 10.64898/2026.08.31.26361802 medRxiv
Top 1%
0.2%
Show abstract

Objectives Emergency department (ED) crowding is recognised as an important public health concern internationally, and is driven principally by exit block, the shortage of inpatient beds for patients requiring admission. This study aimed to evaluate whether a complex intervention targeting hospital occupancy improved ED patient flow, and quantified the change in attendances. Methods A controlled interrupted time series using weekly, publicly reported Public Health Scotland data from 1 January 2022 to 1 February 2026. The multi-component intervention focused on reducing hospital occupancy and included additional adult social care funding; engagement with regional social care providers; accelerated implementation of the Discharge without Delay programme; re-evaluation of whole-hospital escalation thresholds and response; resource and data supporting inpatient department reductions in length of stay; and additional investment in remote clinical assessment. The intervention commenced at a large tertiary ED on 01 February 2025. Primary outcomes were the proportions of attendances spending [&ge;]4, [&ge;]8 and [&ge;]12 hours in the ED. The secondary outcome was attendance volume. Segmented regression was fitted with a contemporaneous control series, seasonal terms and autoregressive moving average errors. Long waits were additionally illustrated as potentially avoided deaths. Results The analysis covered 161 pre-intervention and 52 post-intervention weeks. Relative to pre-intervention levels, the proportion of attendances waiting over 4 hours fell by 10.4% (95% CI 1.6 to 19.2%), by 16.4% (95%CI 1.3 to 31.5%) over 8 hours and by 24.3% (95%CI 2.6 to 46.1%) over 12 hours. Using established associations between long ED waits and excess mortality, by one-year the intervention was potentially associated with 54 fewer excess deaths (95%CI 19 to 93). Attendances rose by 3.8% (95%CI 1.3 to 6.4%) against the counterfactual. Conclusions A complex intervention targeting hospital occupancy was associated with a reduction in long ED waits despite rising attendances. Interventions addressing hospital occupancy can meaningfully improve ED crowding.

11
An interpretable, formally verified point-of-care ultrasound risk equation for difficult videolaryngoscopy: development and internal validation

Oyarzun-Silva, R. A.; Hernandez-Hernandez, P.; Fernandez-Vaquero, M. A.; De Luis-Cabezon, N.

2026-09-02 anesthesia 10.64898/2026.08.28.26361621 medRxiv
Top 1%
0.2%
Show abstract

Background. Videolaryngoscopy still requires adjuncts or hyperangulated rescue in a clinically important minority, and bedside screening discriminates modestly. Point-of-care ultrasound (POCUS) of the anterior airway is a promising alternative, but existing prediction models are opaque or assume a pre-specified functional form. We developed and internally validated a parsimonious, fully disclosed POCUS risk equation whose form is recovered from data and whose structural properties are machine-checked by formal proof - to our knowledge the first formally verified clinical risk predictor - following TRIPOD+AI 2024. Methods. In a prospective single-centre, single-operator cohort of 259 adults undergoing elective videolaryngoscopy (no-Easy airway 68/259, 26.3%), Sequentially Thresholded Least Squares with bootstrap stability selection (B=300) screened a 71-term library of nine POCUS features and retained a seven-term logistic equation; a two-term bootstrap-stable model was pre-specified as robustness analysis. Internal validation used 5x10 repeated cross-validation plus temporal and device hold-outs, with pre-specified overfitting and optimism assessments. Five behavioural properties of the deployed equation were machine-checked in Lean 4. Results. Two interactions met the |c|/sigma_c>2 stability criterion: skin-to-epiglottis x skin-to-hyoid-bone distance and tongue volume x sagittal tongue area. The seven-term equation reached a 5x10 cross-validated C-statistic of 0.966 (optimism-corrected 0.968) and held across temporal and device hold-outs (0.94-0.97). Calibration-in-the-large matched prevalence, with cross-validated slope 0.90 attenuating to 0.625 out-of-time; standard recalibration restored 0.92 without loss of discrimination. The pre-specified two-term robustness model reproduced this performance (C-statistic 0.964-0.968; events-per-parameter 34; shrinkage 0.99), confirming the result is not an artefact of the screening stage. Net benefit over a clinical baseline was positive across 10-50% thresholds. All five Lean 4 theorems compiled without sorry. Conclusions. A sparse, formally verified POCUS equation predicts difficult videolaryngoscopy with high internally validated discrimination and quantified, modest overfitting. Because the equation was developed in a single-operator cohort and its inputs are operator-dependent, external validation requires prior harmonisation of the measurement protocol and operator credentialing.

12
BanffNET, a Deep Learning System for Comprehensive Histological Lesion Quantification in Kidney Transplant Biopsies

Buzzanca, G.; Pala, C.; He, J.; Hofstraat-Boersma, R.; Tammaro, A.; van Midden, D.; Buelow, R.; Hoelscher, D. L.; Muehlfeld, A. S.; Koeller, m.; Kozakowski, N.; Boehmig, G.; Halloran, P. F.; van der Helm, D.; Meziyerh, S.; Venhuizen, J.-H.; Haitjema, S.; Dijkstra, J.; Hilbrands, L. B.; Steenbergen, E. J.; van Zuilen, A. D.; Nurmohamed, A. S.; Bemelman, F. J.; Bruns, I. B.; Callegaro, G.; van de Water, B.; Pieters, T. T.; Breimer, G. E.; Rossi, G. M.; Fiaccadori, E.; Maggiore, U.; Roelofs, J. J. T. H.; Testa, F.; Fontana, F.; Abiola, A. A.; Delsante, M.; Corthals, G. L.; Peters-Sengers, H.; Ngu

2026-09-02 pathology 10.64898/2026.08.28.26360029 medRxiv
Top 1%
0.2%
Show abstract

Accurate, reproducible interpretation of kidney allograft biopsies is critical for diagnosis of graft injury to guide prognosis and management. The international Banff classification is a consensus diagnostic system based on semiquantitative histological lesion scoring on either extent or severity of kidney transplant biopsies. However, pathologist scoring is limited by substantial interobserver variability, constrained scalability, and the inherent nature of the scoring system itself. Here we present BanffNET, a weakly supervised, probabilistic deep learning framework that combines self-supervised feature extraction with a novel Bayesian multiple-instance learning framework to predict (continuously) the full spectrum of Banff lesion scores directly from whole-slide images (WSIs). Using lesion-specific aggregation functions tailored to localized (modeling lesion severity) and diffuse pathologies (modeling lesion extent), BanffNET generates interpretable, patch-level probability maps and calibrated slide-level scores. BanffNET's performance was assessed relative to consensus, biological correlates of rejection and clinical outcome, demonstrating superior consistency, transportability and generalization. Trained on 7,249 WSIs from three cohorts, BanffNET demonstrates consistent performance on 11,028 WSIs across five external test sets, performing on par or exceeding expert consensus across lesions. BanffNET scores align more closely than pathologist Banff scores with molecular profiles of rejection, offering a transparent, biologically grounded framework for computational pathology with relevance beyond transplantation.

13
Acute Renal, Hepatic, Thromboembolic and Functional Complications after Community-Acquired Acute Lower Respiratory Tract Infection: A Prospective Cohort Study in Bristol, UK, 2022-2024

Chatzilena, A.; Hyams, C.; Challen, R.; Lahuerta, M.; McGuinness, S.; Clout, M.; Begier, E.; King, J.; Morales-Aza, B.; Duale, K.; Rodriguez Pereira, A.; Healy, W.; Southern, J.; Wells, P.; Lihou, K.; Grimes, C.; Campling, J. A.; Maskell, N.; Oliver, J.; Vyse, A.; Gessner, B.; Finn, A.; Danon, L.; The AvonCAP Research Group,

2026-09-02 respiratory medicine 10.64898/2026.08.28.26361617 medRxiv
Top 1%
0.2%
Show abstract

Introduction Acute lower respiratory tract disease (aLRTD) is a leading cause of hospitalisation and death, particularly in older adults and adults with comorbidities, with acute lower respiratory tract infection (aLRTI; pneumonia and non-pneumonic LRTI) being a major component. Non-pulmonary complications and functional decline after aLRTI are recognised, but their pathogen-specific burden is poorly described. We aimed to quantify renal, hepatic, thromboembolic and functional complications, and mortality, after aLRTI hospitalisation, by clinical phenotype and pathogen. Methods We conducted a cohort study of adults (>18 years) admitted with aLRTD to two hospitals in Bristol, UK (01 August 2022-31 July 2024). aLRTD was classified as pneumonia, non-pneumonic LRTI (NP-LRTI) or no diagnosis of aLRTI. Pathogens were identified from standard-of-care and research microbiology. Outcomes were acute kidney injury (AKI), acute liver dysfunction, venous thromboembolism (VTE), in-hospital falls, reduced mobility at discharge, increased care requirements, and 30-day and 1-year mortality. Analyses were descriptive. Results Among 246,797 adult admissions, 21,456 aLRTD hospitalisations were included: 10,239 (47.7%) pneumonia, 7,742 (36.1%) NP-LRTI and 3,475 (16.2%) with no evidence of aLRTI. Of 19,152 tested aLRTD admissions, 8,503 (44.4%) had a positive microbiological/virological test, yielding 9,204 pathogen detections; 1,194 (6.2%) had co-infections, and SARS-CoV-2 was most frequent, with influenza the second most common in pneumonia and NP-LRTI. Pneumonia had greater severity than NP-LRTI and no diagnosis of aLRTI (median length of stay 6 vs 4 vs 4 days; ICU admission 3.4% vs 0.7% vs 0.5%, respectively). Overall, 22.2% developed AKI, 6.1% acute liver dysfunction, 0.6% DVT and 2.4% PE; 1.8% had a fall, 11.5% reduced mobility, and 16.6% required increased care at discharge. 30-day and 1-year mortality were highest for pneumonia (14.0% and 32.0%, respectively). Pathogen-specific analyses showed longer stays and higher complications and mortality rates for SARS-CoV-2 and Streptococcus pneumoniae, and shorter stays with lower complication and mortality rates for influenza and Haemophilus influenzae. Conclusions Non-cardiovascular complications and functional decline after aLRTI were common, particularly in pneumonic and SARS-CoV-2 or pneumococcal disease. These findings support routine surveillance for renal, hepatic, thromboembolic events, early mobilisation and rehabilitation, and consideration of multi-system outcomes when evaluating public health and economic value of vaccines and therapies.

14
AI Video Analysis of Psychomotor Performance in EMS Education: Agreement With Human Evaluators Across Three Skills

Otte, J. H.; Cartagena, A.

2026-08-31 medical education 10.64898/2026.08.26.26361437 medRxiv
Top 1%
0.2%
Show abstract

Background. A primary constraint on the capacity of EMS programs to meet industry demand is psychomotor instruction and verification, requiring direct observation of each student by a qualified evaluator. Whether AI video analysis can relieve it is untested; none has been applied to EMS skill examination or compared with human examiners. Objective. To quantify human EMS evaluator inter-rater reliability and evaluate an AI video-analysis platform against it. Methods. In a prospective, fully crossed study, five certified EMS evaluators and an AI platform independently scored identical video-recorded EMT performances of cervical collar application (n=15), bag-valve-mask (BVM) ventilation (n=14), and medical assessment (n=15) on dichotomous checklists with critical-failure criteria. Agreement was assessed at item, score, and decision levels using Fleiss' kappa, Krippendorff's alpha, Gwet's AC1, and ICC(2,1)/ICC(2,k). Results. Human item agreement was moderate (kappa 0.409 to 0.467), as was single-rater reliability (ICC(2,1) 0.539 to 0.694), against good panel reliability (ICC(2,k) 0.854 to 0.919). Recorded pass/fail agreement was fair (kappa 0.297 to 0.388) and critical-failure agreement near zero for two skills (kappa 0.028, 0.119). AI alignment tracked rubric observability rather than task complexity: r = 0.857 (collar, exceeding every human), -0.173 (BVM), 0.664 (medical), and it was most lenient on two skills. Conclusions. Human evaluators are an imperfect standard, especially on critical failures. The AI was a legitimate additional rater where checklist items were discrete and visually verifiable, but not where credit required judging continuous quantities such as ventilation rate, volume, or suction duration. Defensible uses are formative and archival, not summative. These results reflect an early, non-specialist configuration: a baseline, not a limit.

15
Vaginal Steaming Practices Among Women At Mulago Sti Clinic, Uganda: Prevalence, Lactobacilli Proportion Differences, And Associated Factors.

Kakai, D.; Twinamasiko, N.; Kigozi, E.; Namutale, R.; Mutesi, B. A.; Bagaya, J.; Akinyi, L.; Kajumbula, H.; Nakubulwa, S.

2026-09-01 obstetrics and gynecology 10.64898/2026.08.28.26361407 medRxiv
Top 1%
0.2%
Show abstract

Abstract Background: Vaginal steaming has gained popularity among women for reasons known best to them. However, the effects of vaginal steaming on vaginal lactobacilli levels remain poorly understood and understudied. This study investigated the prevalence, assessed differences in the presence of bacterial vaginosis (BV) among women who practiced vaginal steaming and those who do not and determined the factors associated with vaginal steaming among women attending Mulago Sexually Transmitted Infections (STI) clinic in Uganda. Methods: This study utilized a cross-sectional design to enroll 181 women aged 18 to 49 years who were systematically sampled at Mulago STI clinic. Interviews were conducted to obtain the demographic characteristics of the participants. Vaginal swabbing and Gram staining were done to attain lactobacilli counts by microscopy which were categorized using the Nugent score. Data were analyzed using Stata. The potential confounding effects of other variables on the relation between vaginal steaming and the presence of bacterial vaginosis as well as factors associated with vaginal steaming were assessed using modified Poisson regression. Results: Prevalence of vaginal steaming was 40.3%, (95% confidence interval (CI) 33.0% - 48.0%). There were 41.1% women who practiced vaginal steaming occasionally, 53.4% who used hot water having herbs and 78.1% who practiced vaginal steaming for medical reasons. There was no difference in the presence of bacterial vaginosis when women who practiced vaginal steaming were compared to those who did not (p-value = 0.286). Factors that were significantly associated with vaginal steaming included having experienced vaginal issues (aPR = 0.07, 95% CI 0.01 - 0.12, p value = < 0.001) and contraceptive use (aPR= 0.52, 95% CI 0.37 - 0.72, p value = 0.001). Conclusions: About 2 in every 5 women at Mulago STI clinic reported to have indulged in vaginal steaming. There was no difference in the presence of bacterial vaginosis when women who practiced vaginal steaming were compared to those who did not. Having experienced vaginal issues and contraceptive use were significantly associated with vaginal steaming among women at Mulago STI clinic, Uganda. The Ministry of Health of Uganda should establish targeted screening and treatment for bacterial vaginosis alongside other sexually transmitted infections.

16
Certified large language model-based diagnostic decision support in rheumatology: the ALLIANCE multicentre randomised controlled trial

Kremer, P.; Schlicker, N.; Hasnaj, R.; Bamberger, J.; Witte, T.; Haase, I.; Mayr, A.; Schmidt, C.; Osteras, N.; Baraliakos, X.; Kuhn, S.; Krusche, M.; Knitza, J.

2026-09-02 rheumatology 10.64898/2026.08.29.26361715 medRxiv
Top 1%
0.2%
Show abstract

Objectives To evaluate whether access to a certified large language model (LLM)-based clinical decision support system improves physician diagnostic performance in rheumatology compared with conventional diagnostic resources alone. Methods In this multicentre, open-label, randomised controlled trial, 82 physicians from seven hospitals in two countries were randomised 1:1 to conventional diagnostic resources plus Prof. Valmed or conventional resources alone. Participants assessed three rheumatology vignettes before and after assistance. The primary outcome was top-1 diagnostic accuracy. Secondary outcomes included top-3 accuracy, diagnostic reasoning, confidence, case-processing time and perceived support quality. Results Top-1 accuracy increased from 22.2% to 33.3% in the intervention group and from 23.3% to 35.0% in the control group, with no between-group difference in improvement (adjusted OR 0.99, 95% CI 0.45 to 2.19; p=0.979). Differences in top-3 accuracy, diagnostic reasoning and confidence were also not significant. Assisted case-processing time was substantially shorter with LLM support (94 vs 206 s; adjusted mean difference -112 s, 95% CI -141 to -83; p<0.001). Information timeliness and perceived diagnostic support quality were rated significantly higher in the intervention group. Exploratory analyses showed persistent overconfidence and substantial AI over-reliance. Conclusions Certified LLM-based diagnostic support did not improve diagnostic accuracy compared with conventional resources, but substantially reduced case-processing time and improved perceived support quality. These findings suggest potential workflow benefits while highlighting overconfidence and over-reliance as important safety considerations.

17
Evaluation of risk stratification at presentation using the Alinity high-sensitivity cardiac troponin I assay

Li, Z.; Fujisawa, T.; Skadberg, O.; Fineran, P.; Thurston, A. J.; Tew, Y. Y.; Aakre, K. M.; Mills, N. L.; Wereski, R.; the POC-ET Investigators,

2026-08-31 cardiovascular medicine 10.64898/2026.08.29.26361405 medRxiv
Top 1%
0.2%
Show abstract

Background: High-sensitivity cardiac troponin (hs-cTn) assays enable safe early discharge of patients at very low risk for myocardial infarction. We previously developed a single-sample rule-out pathway using the ARCHITECT hs-cTnI assay to risk stratify patients with suspected acute coronary syndrome. In a secondary analysis of the POC-ET (Point of Care Evaluation of High-sensitivity Cardiac Troponin) study, we evaluated performance of risk stratification with the Alinity hs-cTnI assay. Methods: Patients presenting with possible myocardial infarction in the POC-ET (NCT05665127) study were included. The primary outcome was type 1, 4b or 4c myocardial infarction or cardiac death at 30 days. Cardiac troponin I (cTnI) was measured in stored materials using the ARCHITECT and Alinity hs-cTnI assays. The sex-specific 99th percentile upper reference limit (URL) are 34 ng/L in men and 16 ng/L in women for both assays. Agreement was assessed with Bland-and-Altman limit of agreement method, Passing Bablok regression, and Pearson's correlation coefficient. Distributions of presentation measurements were compared with Kolmogorov-Smirnov test. Performance was evaluated in the overall population and prespecified subgroups. The negative predictive value (NPV) and sensitivity were determined and proportion of patients identified as low, intermediate, and high risk were calculated and modelled using ordinal logistic regression. Results: In 986 patients (60 [51-70] years, 38% female), 78 (7.9%) had a primary outcome. Strong agreement was found in the raw cTnI measurements (99% samples within the Bland-Altman limit of agreement; correlation coefficient r: 0.967 (95% CI 0.964-0.969, P<0.001); Passing Bablok regression: slope 1.12 [1.11-1.13], intercept -0.16 [-0.18 to -0.13]). At presentation, distributions of cTnI measurements by the two assays were similar (P=0.810). Both assays showed comparable diagnostic performance using a risk stratification threshold of <5 ng/L and the sex-specific diagnostic threshold, with the same NPV (Alinity 100 [99.7-100]% versus ARCHITECT 100 [99.7-100]%) and sensitivity (Alinity 100 [97.3-100]% versus ARCHITECT 100 [97.3-100]%). Similar proportions of patients stratified as low- (Alinity 67% versus ARCHITECT 67%), intermediate-risk (23% versus 24%) and high-risk (10% versus 9%) at presentation with minor reclassification. Similar efficacy was observed across subgroups stratified by sex, age, history of myocardial infarction, renal function, and symptom duration. Conclusions: The Alinity hs-cTnI and the ARCHITECT hs-cTnI assays can be used interchangeably in the assessment of suspected myocardial infarction with comparable safety and efficacy.

18
Hiatal Hernia Size and De Novo Gastroesophageal Reflux Disease After Sleeve Gastrectomy: A Single-Center Retrospective Study

Ricarte Almeida, E. R.; Mata Quintero, C. J.; Sesma Chazaro, J.; Peralta Rivera, C.; Arteaga Gonzalez, C. D.

2026-09-02 surgery 10.64898/2026.08.31.26361833 medRxiv
Top 1%
0.2%
Show abstract

Background: Sleeve gastrectomy is the most frequently performed bariatric procedure worldwide but is associated with the development of de novo gastroesophageal reflux disease (GERD). Hiatal hernia has been identified as a relevant anatomical factor in postoperative reflux, although most studies evaluate it dichotomously without analyzing whether its size influences GERD risk. The aim was to evaluate the association between preoperative hiatal hernia size and de novo GERD after sleeve gastrectomy. Methods: Retrospective, single - center, observational study of patients undergoing sleeve gastrectomy at Hospital Central Norte de Petroleos Mexicanos (2018 - 2025). Demographic and clinical characteristics, endoscopic classification of hiatal hernia size (small <2 cm, medium 2.1 - 4 cm, large >4 cm), and evidence of de novo GERD were analyzed using descriptive statistics, Fisher's exact test, odds ratio (=R) estimation with 95% confidence intervals (CI), and binary logistic regression. Statistical significance was set at p<0.05. Results: Fiftysix patients were included (mean age 48.3 {+/-} 8.1 years; 67.9% male). Hiatal hernia classification was conclusive in 46 patients (82.1%): 63.0% no hernia, 4.3% small, 30.4% medium, and 2.2% large. De novo GERD occurred in 14.0% of patients without preexisting GERD (6/43). No significant association was found between hiatal hernia size and de novo GERD (Fisher p=0.515). In the reduced logistic model, neither hiatal hernia (medium/large vs. absent/small; OR 3.47; 95% CI 0.50 - 29.43; p=0.207) nor age (OR 1.02; 95% CI 0.90 - 1.13; p=0.754) was significantly associated. No evaluated factor (sex, smoking, alcohol, age) reached significance. Conclusions: In this cohort, no statistically significant association was demonstrated between preoperative hiatal hernia size and de novo GERD after sleeve gastrectomy; however, the low number of events limits the ability to exclude a clinically relevant association. These findings are compatible with a multifactorial mechanism rather than with the isolated presence of this finding. Prospective studies with larger sample sizes and standardized reflux assessment instruments are required to confirm these results.

19
Evaluating Mean Platelet Volume in relation to Disease Severity in Paediatric Sickle Cell Anaemia: A Cross-Sectional Study in Kwara State, North-Central Nigeria

Oladimeji, F. D.; Adewoyin, A. D.; Oyeleke, K. O.

2026-09-02 hematology 10.64898/2026.08.28.26361349 medRxiv
Top 1%
0.1%
Show abstract

Background: Sickle cell anaemia (SCA) is characterised by chronic haemolysis, inflammation, platelet activation, and recurrent vaso-occlusive complications. Mean platelet volume (MPV) is a readily available platelet index, but evidence regarding its relationship with disease severity in paediatric SCA remains limited and inconsistent, particularly in African populations. Objective: To evaluate the relationship between MPV and disease severity among children with SCA in Kwara State, North-Central Nigeria. Methods: This hospital-based cross-sectional study included 51 clinically stable children with confirmed SCA consecutively recruited from the paediatric haematology clinic of Children Emergency Specialist Hospital, Ilorin. Complete blood count, including MPV, was performed using a Rayto RT-7600 automated haematology analyser. Disease severity was assessed using a composite clinical and laboratory scoring system based on a previously described method. Pearson's correlation, Spearman's rank correlation, simple linear regression, and the Kruskal-Wallis test were used as appropriate. Statistical significance was set at p < 0.05. Results: Of 51 participants, 14 (27.5%) had mild, 33 (64.7%) moderate, and 4 (7.8%) severe disease. Mean MPV was 9.34 +/- 0.76 fL (range, 8.0-11.2). Pearson's correlation showed a weak positive, non-significant linear relationship with severity score (r = 0.231, p = 0.103), whereas Spearman's analysis showed a weak positive monotonic association (rho = 0.286, p = 0.042). Regression explained 5.3% of severity-score variation (R2 = 0.053, p = 0.103). MPV did not differ significantly across severity categories (H = 2.163, p = 0.339). MPV correlated inversely with haemoglobin (r = -0.556, p < 0.001) and positively with platelet count (r = 0.307, p = 0.029). Conclusion: MPV showed a weak relationship with disease severity but inconsistent statistical evidence across analyses. The limited explained variance and absence of significant differences between severity categories do not support MPV as a standalone severity marker. Larger longitudinal studies are warranted. Keywords: Sickle cell anaemia; Mean platelet volume; Disease severity; Platelet indices; Paediatric haematology; Cross-sectional study; Nigeria.

20
The dynamics of arterial pressure itself predict intraoperative hypotension beyond its current value: an interpretable additive model validated in 3,069 external patients under a selection-bias-resistant protocol

Oyarzun, R.; Hernandez, P.

2026-08-31 anesthesia 10.64898/2026.08.26.26361468 medRxiv
Top 1%
0.1%
Show abstract

Background. Whether predictors of intraoperative hypotension (IOH) add information beyond the mean arterial pressure (MAP) already displayed on the monitor is contested: selection bias in common evaluation designs inflates apparent performance, and the field has called for comparisons against simple MAP-based references under bias-resistant protocols. Existing predictors also depend on proprietary waveform analysis or pulse-contour monitors, restricting both deployment and external validation. Methods. Using 807 non-cardiac surgery patients from the open VitalDB database, we derived an additive gradient boosting model (one split per tree: a learned shape function per variable, no interactions) from three variables computable from an arterial line alone: current MAP, its drift from the patient's own 20-minute baseline, and the growth of its rolling variance (critical slowing down). Evaluation used patient-level 5-fold cross-validation under a strict protocol - exclusion of the 65-75 mmHg grey zone and of all samples already hypotensive at prediction time - with MAP alone (same learner class) as comparator. The frozen model was then validated, without any refitting, on an independent cohort from another continent (MOVER, University of California Irvine) following a pre-registered plan sealed before external data access. Results. In development the pressure-only model reached AUROC 0.907 vs. 0.884 for MAP alone (Delta AUROC +0.023, 95% CI +0.017 to +0.029) at 5 min, with +0.031 and +0.032 at 10 and 15 min, and good calibration (Brier skill +0.418 vs. prevalence). In external validation on 3,069 patients (442,194 samples, 1-minute charting, event prevalence 5.8%), the advantage not only transferred but was larger than in development: AUROC 0.696 vs. 0.638, Delta AUROC +0.058 (95% CI +0.051 to +0.064), meeting both pre-registered gates. Discrimination transferred; calibration did not (external Brier skill -0.014), requiring local recalibration. In the unrestricted scenario, where samples already at threshold are retained, the advantage collapsed (+0.007), reproducing the selection effect this paper documents. A secondary model adding pulse-contour cardiac output and stroke volume variation improved development discrimination further (Delta AUROC +0.035) but could be externally validated in only 39 patients, because those signals are rarely recorded. Conclusions. The dynamics of arterial pressure itself - drift from a patient-specific baseline and variance growth - carry predictive information beyond its current value, in a fully interpretable additive model that requires only an arterial line, no waveform access and no proprietary hardware. The advantage is confirmed in a pre-registered frozen-model external validation of over three thousand patients, and is largest at coarse recording cadence, where instantaneous pressure is least informative.